NSF PAR Search | NSF Public Access Repository

Note: When clicking on a Digital Object Identifier (DOI) number, you will be taken to an external site maintained by the publisher. Some full text articles may not yet be available without a charge during the embargo (administrative interval).
What is a DOI Number?

Some links on this page may take you to non-federal websites. Their policies may differ from this site.

Active Human Pose Estimation via an Autonomous UAV Agent

https://doi.org/10.1109/IROS58592.2024.10801780

Chen, Jingxi; He, Botao; Singh, Chahat Deep; Fermüller, Cornelia; Aloimonos, Yiannis (October 2024, IEEE)

Full Text Available
Minimal perception: enabling autonomy in resource-constrained robots

https://doi.org/10.3389/frobt.2024.1431826

Singh, Chahat Deep; He, Botao; Fermüller, Cornelia; Metzler, Christopher; Aloimonos, Yiannis (September 2024, Frontiers in Robotics and AI)

The rapidly increasing capabilities of autonomous mobile robots promise to make them ubiquitous in the coming decade. These robots will continue to enhance efficiency and safety in novel applications such as disaster management, environmental monitoring, bridge inspection, and agricultural inspection. To operate autonomously without constant human intervention, even in remote or hazardous areas, robots must sense, process, and interpret environmental data using only onboard sensing and computation. This capability is made possible by advancements in perception algorithms, allowing these robots to rely primarily on their perception capabilities for navigation tasks. However, tiny robot autonomy is hindered mainly by sensors, memory, and computing due to size, area, weight, and power constraints. The bottleneck in these robots lies in the real-time perception in resource-constrained robots. To enable autonomy in robots of sizes that are less than 100 mm in body length, we draw inspiration from tiny organisms such as insects and hummingbirds, known for their sophisticated perception, navigation, and survival abilities despite their minimal sensor and neural system. This work aims to provide insights into designing a compact and efficient minimal perception framework for tiny autonomous robots from higher cognitive to lower sensor levels.
more » « less
Full Text Available
AcTExplore: Active Tactile Exploration on Unknown Objects

https://doi.org/10.1109/ICRA57147.2024.10611667

Shahidzadeh, Amir-Hossein; Yoo, Seong Jong; Mantripragada, Pavan; Singh, Chahat Deep; Fermüller, Cornelia; Aloimonos, Yiannis (May 2024, IEEE)

Full Text Available
Microsaccade-inspired event camera for robotics

https://doi.org/10.1126/scirobotics.adj8124

He, Botao; Wang, Ze; Zhou, Yuan; Chen, Jingxi; Singh, Chahat Deep; Li, Haojia; Gao, Yuman; Shen, Shaojie; Wang, Kaiwei; Cao, Yanjun; et al (May 2024, Science Robotics)

Neuromorphic vision sensors or event cameras have made the visual perception of extremely low reaction time possible, opening new avenues for high-dynamic robotics applications. These event cameras’ output is dependent on both motion and texture. However, the event camera fails to capture object edges that are parallel to the camera motion. This is a problem intrinsic to the sensor and therefore challenging to solve algorithmically. Human vision deals with perceptual fading using the active mechanism of small involuntary eye movements, the most prominent ones called microsaccades. By moving the eyes constantly and slightly during fixation, microsaccades can substantially maintain texture stability and persistence. Inspired by microsaccades, we designed an event-based perception system capable of simultaneously maintaining low reaction time and stable texture. In this design, a rotating wedge prism was mounted in front of the aperture of an event camera to redirect light and trigger events. The geometrical optics of the rotating wedge prism allows for algorithmic compensation of the additional rotational motion, resulting in a stable texture appearance and high informational output independent of external motion. The hardware device and software solution are integrated into a system, which we call artificial microsaccade–enhanced event camera (AMI-EV). Benchmark comparisons validated the superior data quality of AMI-EV recordings in scenarios where both standard cameras and event cameras fail to deliver. Various real-world experiments demonstrated the potential of the system to facilitate robotics perception both for low-level and high-level vision tasks.
more » « less
Full Text Available
Ajna: Generalized deep uncertainty for minimal perception on parsimonious robots

https://doi.org/10.1126/scirobotics.add5139

Sanket, Nitin J.; Singh, Chahat Deep; Fermüller, Cornelia; Aloimonos, Yiannis (August 2023, Science Robotics)

Robots are active agents that operate in dynamic scenarios with noisy sensors. Predictions based on these noisy sensor measurements often lead to errors and can be unreliable. To this end, roboticists have used fusion methods using multiple observations. Lately, neural networks have dominated the accuracy charts for perception-driven predictions for robotic decision-making and often lack uncertainty metrics associated with the predictions. Here, we present a mathematical formulation to obtain the heteroscedastic aleatoric uncertainty of any arbitrary distribution without prior knowledge about the data. The approach has no prior assumptions about the prediction labels and is agnostic to network architecture. Furthermore, our class of networks, Ajna, adds minimal computation and requires only a small change to the loss function while training neural networks to obtain uncertainty of predictions, enabling real-time operation even on resource-constrained robots. In addition, we study the informational cues present in the uncertainties of predicted values and their utility in the unification of common robotics problems. In particular, we present an approach to dodge dynamic obstacles, navigate through a cluttered scene, fly through unknown gaps, and segment an object pile, without computing depth but rather using the uncertainties of optical flow obtained from a monocular camera with onboard sensing and computation. We successfully evaluate and demonstrate the proposed Ajna network on four aforementioned common robotics and computer vision tasks and show comparable results to methods directly using depth. Our work demonstrates a generalized deep uncertainty method and demonstrates its utilization in robotics applications.
more » « less
Full Text Available
NudgeSeg: Zero-Shot Object Segmentation by Repeated Physical Interaction

https://doi.org/10.1109/IROS51168.2021.9635832

Singh, Chahat Deep; Sanket, Nitin J.; Parameshwara, Chethan M.; Fermuller, Cornelia; Aloimonos, Yiannis (September 2021, 2021 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS))

Recent advances in object segmentation have demonstrated that deep neural networks excel at object segmentation for specific classes in color and depth images. However, their performance is dictated by the number of classes and objects used for training, thereby hindering generalization to never seen objects or zero-shot samples. To exacerbate the problem further, object segmentation using image frames rely on recognition and pattern matching cues. Instead, we utilize the ‘active’ nature of a robot and their ability to ‘interact’ with the environment to induce additional geometric constraints for segmenting zero-shot samples. In this paper, we present the first framework to segment unknown objects in a cluttered scene by repeatedly ‘nudging’ at the objects and moving them to obtain additional motion cues at every step using only a monochrome monocular camera. We call our framework NudgeSeg. These motion cues are used to refine the segmentation masks. We successfully test our approach to segment novel objects in various cluttered scenes and provide an extensive study with image and motion segmentation methods. We show an impressive average detection rate of over 86% on zero-shot objects.
more » « less
Full Text Available
0-MMS: Zero-Shot Multi-Motion Segmentation With A Monocular Event Camera

https://doi.org/10.1109/ICRA48506.2021.9561755

Parameshwara, Chethan M.; Sanket, Nitin J.; Singh, Chahat Deep; Fermuller, Cornelia; Aloimonos, Yiannis (January 2021, 2021 IEEE International Conference on Robotics and Automation (ICRA))

Segmentation of moving objects in dynamic scenes is a key process in scene understanding for navigation tasks. Classical cameras suffer from motion blur in such scenarios rendering them effete. On the contrary, event cameras, because of their high temporal resolution and lack of motion blur, are tailor-made for this problem. We present an approach for monocular multi-motion segmentation, which combines bottom-up feature tracking and top-down motion compensation into a unified pipeline, which is the first of its kind to our knowledge. Using the events within a time-interval, our method segments the scene into multiple motions by splitting and merging. We further speed up our method by using the concept of motion propagation and cluster keyslices.The approach was successfully evaluated on both challenging real-world and synthetic scenarios from the EV-IMO, EED, and MOD datasets and outperformed the state-of-the-art detection rate by 12%, achieving a new state-of-the-art average detection rate of 81.06%, 94.2% and 82.35% on the aforementioned datasets. To enable further research and systematic evaluation of multi-motion segmentation, we present and open-source a new dataset/benchmark called MOD++, which includes challenging sequences and extensive data stratification in-terms of camera and object motion, velocity magnitudes, direction, and rotational speeds.
more » « less
Full Text Available

Search for: All records